Papers with video description generation

1 papers
Incorporating Semantic Attention in Video Description Generation (L18-1)

Copied to clipboard

Challenge: Existing methods to generate video descriptions fail to mention objects and actions in videos .
Approach: They propose an LSTM-based sequence-to-sequence model with semantic attention mechanism for video description generation that includes external fine-grained visual information.
Outcome: The proposed model can selectively focus on external fine-grained visual information and have a better quality of video descriptions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations